Papers by Linea Flansmose Mikkelsen
DDisCo: A Discourse Coherence Dataset for Danish (2022.lrec-1)
Copied to clipboard
| Challenge: | Discourse coherence models have been developed using randomly shuffled texts instead of highly edited and coherent data. |
| Approach: | They propose to annotate Danish Wikipedia and Reddit for discourse coherence using real-world text instead of artificially incoherent text for training and testing models. |
| Outcome: | The proposed model performs well on annotated texts from the Danish Wikipedia and Reddit dataset. |